Speech SynthesisTTS이어서 읽기2024.10.02VALL-E 2: Neural Codec Language Models are Human Parity Zero-Shot Text to Speech Synthesizers ↗2024.09.30VALL-E R: Robust and Efficient Zero-Shot Text-to-Speech Synthesis via Monotonic Alignment ↗2024.09.23Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech ↗댓글 보기 · Disqus
댓글 보기 · Disqus